Graphwalks parents >128k
Progress Over Time
Interactive timeline showing model performance evolution on Graphwalks parents >128k
Graphwalks parents >128k Leaderboard
| Context | Cost | License | ||||
|---|---|---|---|---|---|---|
| 1 | Anthropic | — | 1.0M | $5.00 / $25.00 | ||
| 2 | Anthropic | — | 1.0M | $5.00 / $25.00 | ||
| 3 | OpenAI | — | 1.1M | $5.00 / $30.00 | ||
| 4 | OpenAI | — | 1.0M | $2.50 / $15.00 | ||
| 5 | OpenAI | — | 1.0M | $2.00 / $8.00 | ||
| 6 | OpenAI | — | 1.0M | $0.40 / $1.60 | ||
| 7 | OpenAI | — | 1.0M | $0.10 / $0.40 |
What is Graphwalks parents >128k?
A graph reasoning benchmark that evaluates language models' ability to find parent nodes in graphs with context length over 128k tokens, testing long-context reasoning and graph structure understanding.
Graphwalks parents >128k is a text benchmark evaluating models on long context, reasoning, and spatial reasoning tasks. LLM Stats tracks 7 models on this benchmark, scored on a 0–1 scale. The current average is 0.4, with the leader at 1.0.
Compare leaders on the best AI for long context, best AI for reasoning and best AI for spatial reasoning leaderboards.
Current leaders
Claude Opus 4.6 from Anthropic currently leads the Graphwalks parents >128k leaderboard with a score of 0.954 across 7 evaluated AI models.
FAQ
Common questions about the Graphwalks parents >128k benchmark and leaderboard.